Black Forest Labs releases Flux3: The first native multimodal foundation model for generating audio 20 seconds of synchronized audio-visual content at a time
Black Forest Lab releases Flux3, a multimodal model built on Self-Flow architecture, integrating image, video, audio, and motion codecs for unified understanding and generation of physical and digital worlds. It is the first to support native audio generation, producing 20-second synced audio-video clips in one pass. Capabilities include text, image, and video-to-video tasks.....